Speech Emotion Recognition Using Iterative Clustering Technique
نویسنده
چکیده
This paper proposes a method to recognize the emotion present in the speech signal using Iterative clustering technique. We propose Mel Frequency Perceptual Linear Predictive Cepstrum (MFPLPC) as a feature for recognizing the emotions. This feature is extracted from the speech and the clustering models are generated for each emotion. For the Speaker Independent classification technique, preprocessing is done on test speeches and features are extracted. In Kmeans clustering algorithm, the classification is based on the minimum distance between each test vector and centroid of clusters. Mean of the minimum distances for each speech is found out. The test speech belongs to the model which has minimum of averages corresponds to particular emotion. We obtained better recognition rate by using MFPLPC feature for a cluster size 1024. The results are obtained for SAVEE database using data of seven emotional states such as anger, disgust, fear, happy, sad, surprise and neutral. KeyWordsEmotion recognition (ER), Mel Frequency Perceptual Linear Predictive Cepstrum (MFPLPC), Vector Quantization (VQ).
منابع مشابه
Classification of emotional speech using spectral pattern features
Speech Emotion Recognition (SER) is a new and challenging research area with a wide range of applications in man-machine interactions. The aim of a SER system is to recognize human emotion by analyzing the acoustics of speech sound. In this study, we propose Spectral Pattern features (SPs) and Harmonic Energy features (HEs) for emotion recognition. These features extracted from the spectrogram ...
متن کاملA Database for Automatic Persian Speech Emotion Recognition: Collection, Processing and Evaluation
Abstract Recent developments in robotics automation have motivated researchers to improve the efficiency of interactive systems by making a natural man-machine interaction. Since speech is the most popular method of communication, recognizing human emotions from speech signal becomes a challenging research topic known as Speech Emotion Recognition (SER). In this study, we propose a Persian em...
متن کاملSpeech Emotion Recognition Based on Power Normalized Cepstral Coefficients in Noisy Conditions
Automatic recognition of speech emotional states in noisy conditions has become an important research topic in the emotional speech recognition area, in recent years. This paper considers the recognition of emotional states via speech in real environments. For this task, we employ the power normalized cepstral coefficients (PNCC) in a speech emotion recognition system. We investigate its perfor...
متن کاملSpeech Emotion Recognition Using Scalogram Based Deep Structure
Speech Emotion Recognition (SER) is an important part of speech-based Human-Computer Interface (HCI) applications. Previous SER methods rely on the extraction of features and training an appropriate classifier. However, most of those features can be affected by emotionally irrelevant factors such as gender, speaking styles and environment. Here, an SER method has been proposed based on a concat...
متن کاملImproving of Feature Selection in Speech Emotion Recognition Based-on Hybrid Evolutionary Algorithms
One of the important issues in speech emotion recognizing is selecting of appropriate feature sets in order to improve the detection rate and classification accuracy. In last studies researchers tried to select the appropriate features for classification by using the selecting and reducing the space of features methods, such as the Fisher and PCA. In this research, a hybrid evolutionary algorit...
متن کاملذخیره در منابع من
با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید
عنوان ژورنال:
دوره شماره
صفحات -
تاریخ انتشار 2015